OpenAI Jalapeño Chip: What It Means for API Costs
OpenAI unveiled Jalapeño, a custom inference ASIC co-designed with Broadcom and manufactured by TSMC, aiming for 50% lower inference cost per token than Nvidia GPUs. Production is planned for 2027, wi…